NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

How Much Data Is Sufficient to Learn High-Performing Algorithms?

https://doi.org/10.1145/3676278

Balcan, Maria-Florina; Deblasio, Dan; Dick, Travis; Kingsford, Carl; Sandholm, Tuomas; Vitercik, Ellen (October 2024, Journal of the ACM)

Algorithms often have tunable parameters that impact performance metrics such as runtime and solution quality. For many algorithms used in practice, no parameter settings admit meaningful worst-case bounds, so the parameters are made available for the user to tune. Alternatively, parameters may be tuned implicitly within the proof of a worst-case approximation ratio or runtime bound. Worst-case instances, however, may be rare or nonexistent in practice. A growing body of research has demonstrated that a data-driven approach to parameter tuning can lead to significant improvements in performance. This approach uses atraining setof problem instances sampled from an unknown, application-specific distribution and returns a parameter setting with strong average performance on the training set. We provide techniques for derivinggeneralization guaranteesthat bound the difference between the algorithm’s average performance over the training set and its expected performance on the unknown distribution. Our results apply no matter how the parameters are tuned, be it via an automated or manual approach. The challenge is that for many types of algorithms, performance is a volatile function of the parameters: slightly perturbing the parameters can cause a large change in behavior. Prior research [e.g.,12,16,20,62] has proved generalization bounds by employing case-by-case analyses of greedy algorithms, clustering algorithms, integer programming algorithms, and selling mechanisms. We streamline these analyses with a general theorem that applies whenever an algorithm’s performance is a piecewise-constant, piecewise-linear, or—more generally—piecewise-structuredfunction of its parameters. Our results, which are tight up to logarithmic factors in the worst case, also imply novel bounds for configuring dynamic programming algorithms from computational biology.
more » « less
Full Text Available
Learning to Branch: Generalization Guarantees and Limits of Data-Independent Discretization

https://doi.org/10.1145/3637840

Balcan, Maria-Florina; Dick, Travis; Sandholm, Tuomas; Vitercik, Ellen (April 2024, Journal of the ACM)

Tree search algorithms, such as branch-and-bound, are the most widely used tools for solving combinatorial and non-convex problems. For example, they are the foremost method for solving (mixed) integer programs and constraint satisfaction problems. Tree search algorithms come with a variety of tunable parameters that are notoriously challenging to tune by hand. A growing body of research has demonstrated the power of using a data-driven approach to automatically optimize the parameters of tree search algorithms. These techniques use atraining setof integer programs sampled from an application-specific instance distribution to find a parameter setting that has strong average performance over the training set. However, with too few samples, a parameter setting may have strong average performance on the training set but poor expected performance on future integer programs from the same application. Our main contribution is to provide the firstsample complexity guaranteesfor tree search parameter tuning. These guarantees bound the number of samples sufficient to ensure that the average performance of tree search over the samples nearly matches its future expected performance on the unknown instance distribution. In particular, the parameters we analyze weightscoring rulesused for variable selection. Proving these guarantees is challenging because tree size is a volatile function of these parameters: we prove that, for any discretization (uniform or not) of the parameter space, there exists a distribution over integer programs such that every parameter setting in the discretization results in a tree with exponential expected size, yet there exist parameter settings between the discretized points that result in trees of constant size. In addition, we provide data-dependent guarantees that depend on the volatility of these tree-size functions: our guarantees improve if the tree-size functions can be well approximated by simpler functions. Finally, via experiments, we illustrate that learning an optimal weighting of scoring rules reduces tree size.
more » « less
Full Text Available
Leveraging Reviews: Learning to Price with Buyer and Seller Uncertainty

https://doi.org/10.1145/3580507.3597663

Guo, Wenshuo; Haghtalab, Nika; Kandasamy, Kirthevasan; Vitercik, Ellen (July 2023, ACM)
Improved Sample Complexity Bounds for Branch-And-Cut

Balcan, Maria-Florina; Prasad, Siddharth; Sandholm, Tuomas; Vitercik, Ellen (January 2022, 28th International Conference on Principles and Practice of Constraint Programming (CP 2022))

Full Text Available
Structural Analysis of Branch-and-Cut and the Learnability of Gomory Mixed Integer Cuts

Balcan, Maria Florina; Prasad, Siddharth; Sandholm, Tuomas; Vitercik, Ellen (January 2022, Advances in Neural Information Processing Systems)

Full Text Available
Estimating Approximate Incentive Compatibility

Vitercik, Ellen; Balcan, Maria-Florina; Sandholm, Tuomas (January 2019, ACM Conference on Economics and Computation)

Full Text Available
Synchronization Strings: Channel Simulations and Interactive Coding for Insertions and Deletions

https://doi.org/10.4230/LIPIcs.ICALP.2018.75

Haeupler, Bernhard; Shahrasbi, Amirbehshad; Vitercik, Ellen (January 2018, International Colloquium on Automata, Languages and Programming)

Full Text Available
A General Theory of Sample Complexity for Multi-Item Profit Maximization

https://doi.org/10.1145/3219166.3219217

Balcan, Maria-Florina; Sandholm, Tuomas; Vitercik, Ellen (January 2018, EC)

Full Text Available

Search for: All records